Restricted Domain Malay Speech Synthesizer Using Syntax-Prosody Representation

نویسندگان

Sabrina Tiun

Rosni Abdullah

چکیده

The speech synthesis approach required in restricted domain speech application is a synthesizer that has high quality like the speech output of ‘slot-filler’ approach but have at least the least flexibility of the ‘genuine’ speech synthesizer. Thus, in this research study, we propose an alternative approach of creating a speech synthesizer to be used in a restricted domain speech application. In our approach, we use word unit as the primary unit and our speech corpus is represented by syntax-prosody tree structures. Speech synthesis is performed by constructing a syntax-prosody tree of a target input sentence. The construction of the tree is by done by adapting an examplebased syntactic parsing approach and the concatenated of synthesis units from the constructed tree nodes will be the synthesized utterance. For evaluation, we performed MOS subjective evaluation on our speech synthesizer with natural speech and two other Malay TTS system. Based on an ANOVA and T-Tests analysis, we found the overall MOS scores of our speech synthesizer output, sound B was (mean = 3.34, sd = 1.10), the other two Malay TTS system; C (mean = 1.95, sd = 0.72) and D (mean = 1.80, sd = 1.04) and the natural speech, A (mean = 4.71, sd = 0.21). We conclude that our Malay speech synthesizer sounded more natural, easier to listen, more pleasant and more fluent compared to the sounds of the other two Malay TTS systems. As expected, the recorded speech was perceived more natural than the output of our Malay speech synthesizer.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Integrating rule and template-based approaches for emotional Malay speech synthesis

The manipulation of prosody, including pitch, duration and intensity, is one of the leading approaches in synthesizing emotion. This paper reports work on the development of a Malay Emotional synthesizer capable of expressing four basic emotions, namely happiness, anger, sadness and fear for any form of text input with various intonation patterns using the prosody manipulation principle. The sy...

متن کامل

Study on Unit-Selection and Statistical Parametric Speech Synthesis Techniques

One of the interesting topics on multimedia domain is concerned with empowering computer in order to speech production. Speech synthesis is granting human abilities to the computer for speech production. Data-based approach and process-based approach are the two main approaches on speech synthesis. Each approach has its varied challenges. Unit-selection speech synthesis and statistical parametr...

متن کامل

Synthesis of Spoken Messages from Semantic Representations. Semantic-Representation-to-Speech System

A semantic-representation-to-speech system communicates orally the information given in a semantic representation. Such a system must Integrate a text generation module, a phonetic conversion module, a prosodic module and a speech synthesizer We wil l see how the syntactic information elaborated by the text generatlon module is used for both phonetic conversion and prosody, so as to produce the...

متن کامل

Linguistic features weighting for a text-to-speech system without prosody model

This paper presents a Non-Uniform Units selection-based TextTo-Speech synthesizer. Nowadays, systems use prosodic models that do not allow the prosody to vary as far as we should hope, involving a listening comfort degradation. Our system has the advantage to avoid the using of prosodic model. Speech units selection builds its features set exclusively from the linguistic information generated b...

متن کامل

Low footprint High intelligibility Malay speech synthesizer based on Statistical Data

Speech synthesis plays a pivotal role nowadays. It can be found in various daily applications such as in mobile phones, navigation systems, languages learning software and so on. In this study, a Malay language speech synthesizer was designed using hidden Markov model to improve the performance of current Malay speech synthesizer and also extend Malay speech technology. Statistical parametric m...

متن کامل

ذخیره در منابع من

ذخیره در منابع من قبلا به منابع من ذحیره شده

{@ msg_add @}

با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره شماره

صفحات -

تاریخ انتشار 2012

Restricted Domain Malay Speech Synthesizer Using Syntax-Prosody Representation

نویسندگان

چکیده

منابع مشابه

Integrating rule and template-based approaches for emotional Malay speech synthesis

Study on Unit-Selection and Statistical Parametric Speech Synthesis Techniques

Synthesis of Spoken Messages from Semantic Representations. Semantic-Representation-to-Speech System

Linguistic features weighting for a text-to-speech system without prosody model

Low footprint High intelligibility Malay speech synthesizer based on Statistical Data

عنوان ژورنال:

اشتراک گذاری